Papers with large-scale multilingual re-evaluation of PII leakage
Do LLMs Really Memorize Personally Identifiable Information? Revisiting PII Leakage with a Cue-Controlled Memorization Framework (2026.acl-long)
Copied to clipboard
| Challenge: | Large Language Models (LLMs) have been reported to “leak” Personally Identifiable Information (PII) successful PII reconstruction often interpreted as evidence of memorization. |
| Approach: | They propose a principled revision of memorization evaluation for Large Language Models . they propose PII leakage should be evaluated under low lexical cue conditions . |
| Outcome: | The proposed method is based on a multilingual re-evaluation of PII leakage across 32 languages and multiple memorization paradigms. |